Back

Clinical and Translational Science

Wiley

Preprints posted in the last 7 days, ranked by how well they match Clinical and Translational Science's content profile, based on 22 papers previously published here. The average preprint has a 0.03% match score for this journal, so anything above that is already an above-average fit.

1
Bayesian Borrowing of External Information in Clinical Trials: A Comparison of MAP, RMAP, and SAM Priors

Choi, L.; McNeer, E.; Beck, C. A.; Neul, J. L.

2026-08-31 pharmacology and therapeutics 10.64898/2026.08.26.26360843 medRxiv
Top 0.1%
8.1%
Show abstract

Bayesian borrowing of external information can improve trial efficiency, particularly in pediatric and rare disease settings where patient populations are limited, but may introduce bias and inflate the Type~I error rate when the trial differs from external studies. Recent U.S. Food and Drug Administration (FDA) draft Bayesian guidance emphasizes careful evaluation of external information, prior specification, and assessment of operating characteristics. This paper compares three meta-analytic-predictive (MAP)-based methods for Bayesian borrowing: the MAP prior, robust MAP (RMAP) prior, and self-adapting mixture (SAM) prior. An adaptive platform trial design in Rett syndrome is used as a case study. Simulation studies evaluate frequentist operating characteristics under varying prior--data conflict, between-study heterogeneity, treatment effects, and clinically significant differences (CSDs) for the SAM prior. The MAP prior achieved the greatest efficiency when external and current data were compatible but exhibited the largest bias under substantial prior--data conflict. The RMAP priors improved robustness through fixed robust-component weights, whereas the SAM prior adaptively adjusted borrowing and was less sensitive to prior--data conflict while retaining efficiency gains when the data were compatible. Although the CSD influenced the degree of adaptive borrowing, as reflected by effective sample size, it had only a modest impact on frequentist operating characteristics. Sensitivity analyses using a skeptical robust component yielded similar qualitative conclusions, while accentuating the differences between the MAP and RMAP priors. These findings provide guidance for evaluating and selecting MAP-based borrowing strategies before trial implementation, particularly in rare disease settings, consistent with current FDA recommendations.

2
GLP-1 Receptor Agonist Initiation and Anti-VEGF Treatment Frequency in Diabetic Macular Edema: an IRIS(R) Registry Cohort Study

Nagalamadaka, P.; Ross, C. J.; Gilbert, J. B.; Stillman, H.; Ghauri, S. Y.; Dutton, S. M.; Kearney, W.; Li, J. H.; Leong, A.; Singh, R. P.; Krzystolik, M. G.

2026-08-31 ophthalmology 10.64898/2026.08.29.26361426 medRxiv
Top 0.1%
7.5%
Show abstract

Purpose: To evaluate whether initiation of GLP-1 receptor agonists (GLP-1RAs) is associated with anti-VEGF treatment burden in type 2 diabetes patients with diabetic macular edema (DME) in the IRIS(R) Registry (Intelligent Research in Sight). Methods: Incident GLP-1RA initiators were matched 1:1 with controls via Mahalanobis distance matching (9,896 pairs; N=19,792) on sociodemographics, DME risk factors, and factors influencing GLP-1RA prescription including hypertension, obesity, chronic kidney disease. A longitudinal mixed-effects event-study model evaluated monthly anti-VEGF injection frequency over a 36-month window (12 months before through 24 months after initiation), adjusting for DME duration. Visual acuity (VA) and central subfield thickness (CST) were secondary outcomes. Results: Following GLP-1RA initiation, anti-VEGF injection trajectories did not significantly differ between the matched GLP-1RA and control cohorts (interaction coefficients -0.18 to 1.59, P>0.05). Likewise, no differences in VA were observed between cohorts (-0.05 to 0.04 logMAR, P>0.05) or CST (-14.12 to 33.58 {micro}m, P>0.05). Conclusion: In these matched cohorts, GLP-1RA initiation was not associated with the trajectory of anti-VEGF use or changes in VA or CST. Precis We used the American Academy of Ophthalmology IRIS(R) Registry (Intelligent Research in Sight) to identify patients with DME. In 19,792 matched patients, there was no significant reduction in injection frequency post GLP1-RA initiation and no significant change in VA or CST.

3
Prospective In-silico Simulation of the VESALIUS-CV Trial Using Biomedical Knowledge Graph and Real-World Data-Driven AI Modeling

Perlman, A.; Goldstein, N.; Goldman, M.; Shapiro, M.; Barash, E.; Bar, A.; Raveh, T.; Tordjman, E.; Schussheim, H.; Dormont, F.; Matalon, O.

2026-08-31 cardiovascular medicine 10.64898/2026.08.26.26361436 medRxiv
Top 0.2%
3.1%
Show abstract

Background. Cardiovascular-outcomes trials are lengthy, costly, and associated with substantial uncertainty prior to readout. In-silico trial simulation using real-world data (RWD) has emerged as a potential tool to support earlier decision-making; however, evidence of prospective predictive validity, generated prior to trial result disclosure, remains limited. Methods. We applied a semi-mechanistic machine learning framework integrating real-world patient data with biologically informed drug representations to prospectively simulate the VESALIUS-CV trial evaluating evolocumab versus placebo. The simulation model was trained on a combination of patient-level real-world data and a drug-centric knowledge graph and validated for both patient-level and trial-level retrospective predictive performance. The model was then used to simulate VESALIUS-CV before public disclosure of trial results, using a locked model and prespecified eligibility criteria and primary endpoint aligned with the clinical protocol. A patient-level time-to-event model was used to generate virtual trial arms, from which cumulative incidence curves, hazard ratios, confidence intervals, and p-values for major adverse cardiovascular events (MACE) were estimated. Results. In retrospective validation, the model demonstrated strong patient-level discrimination, with time-dependent ROC-AUC values ranging from 0.80 to 0.90 across follow-up horizons. For trial-level validation, 22 randomized cardiovascular-outcomes trials were simulated, and hazard ratios for 3-point MACE across 24 between-arm comparisons showed consistent directional agreement and quantitative correlation with published results such that the model accurately predicted trial success, achieving an F1 score of 0.83, with precision of 0.79 and sensitivity of 0.89. In a fully prospective application, the simulation predicted a statistically significant reduction in 3-point MACE with evolocumab versus placebo, estimating a hazard ratio of 0.78 (95% CI, 0.70-0.87) at 54 months. These predictions were consistent with the subsequently reported VESALIUS-CV results, which demonstrated a hazard ratio of 0.75 (95% CI, 0.65-0.86) at 55 months of median follow-up. Conclusions. In a fully prospective setting, a RWD-driven, AI-based simulation accurately predicted the direction, magnitude, and temporal dynamics of treatment effects observed in the VESALIUS-CV trial. These results demonstrate that in-silico trial simulation can anticipate clinical outcomes in the prospective setting, supporting its use as a complementary tool for early decision-making, trial design optimization, and de-risking in cardiovascular drug development.

4
Positive end-expiratory pressure versus sham valve/zero end-expiratory pressure in cardiopulmonary resuscitation during manual ventilation toimprove neurological outcomes in adult patients suffering an out-of-hospital cardiac arrest - an investigator-initiated, pragmatic, registry-based, multicenter, parallel-group, triple-blind randomized controlled superiority clinical trial in the ARREST registry (REVIVE-PEEP protocol Stage-1 Registered Report)

van Eijk, J.; Schober, P.; van Schuppen, H.; ter Schure, J.

2026-08-31 emergency medicine 10.64898/2026.08.27.26361533 medRxiv
Top 0.2%
2.0%
Show abstract

We present our Stage-1 Registered Report as a full clinical trial article with all methods in past tense and including mock results, table and figures for the primary analysis. To remind the reader that this Stage-1 article is written before data collection, we highlight in color that these mock results are only for illustrative purposes and will be replaced by the actual results in the Stage-2 Registered Report. Background In patients experiencing out-of-hospital cardiac arrest, optimization of oxygen delivery during cardiopulmonary resuscitation is a critical. Although both positive end-expiratory pressure (PEEP) and zero end-expiratory pressure (ZEEP) are employed during CPR, their respective impacts on clinically relevant outcomes is yet to be clearly established. Methods This investigator-initiated, pragmatic, registry-based, multicenter, triple-blind randomized controlled superiority trial evaluates whether applying 8 cm H2O PEEP during cardiopulmonary resuscitation improves outcomes compared with ZEEP in adults with non-traumatic, non-drowning out-of-hospital cardiac arrest. Pre-randomized CPR kits (1:1 PEEP vs. sham) were used by ambulance sites during manual ventilation throughout the resuscitation process. The primary analysis was conducted in the principal stratum of patients who received either a supraglottic airway or endotracheal tube. The primary outcome was neurological status at hospital discharge measured by a utility-weighted score on the modified Rankin Scale. Secondary outcomes included prehospital return of spontaneous circulation, 30-day survival, and 6-month quality of life. The primary safety outcome was clinically significant pneumothorax.

5
Spectral and melanopic dose calibration of consumer see-through extended-reality glasses for controlled retinal photostimulation

Gaidica, M.; Rosengart, M.

2026-08-31 ophthalmology 10.64898/2026.08.26.26361398 medRxiv
Top 0.3%
1.7%
Show abstract

Light reaching the retina is a primary regulator of human circadian physiology, acting largely through melanopsin-expressing retinal ganglion cells with peak short-wavelength sensitivity. Delivering known, repeatable retinal doses outside the laboratory is difficult because conventional light sources leave viewing geometry, gaze, and ambient conditions uncontrolled. Consumer extended-reality (XR) glasses fix a bright binocular display in constant geometry relative to the eye, but their suitability as calibrated photic stimulators has not been established. Here we validate a commercial micro-OLED XR display (VITURE Luma Ultra) for controlled retinal photostimulation. A purpose-built host application renders exact 8-bit RGB stimuli while independently controlling hardware brightness and logging all intensity-determining state; spectral radiance was measured at the retinal position of a 3D-printed phantom head with an open-source miniature spectroradiometer, anchored to absolute units by a luminance transfer calibration. The blue primary peaks at 461 nm (FWHM 43 nm), is spectrally invariant across a >10-fold intensity range, and at maximum output delivers an estimated 299 lx melanopic equivalent daylight illuminance, above consensus daytime recommendations, while remaining roughly two orders of magnitude below photobiological safety limits. The red primary is visually effective with minimal melanopic drive (melanopic DER 0.10), enabling spectrally shifted evening stimulation. Unlike the immersive virtual-reality headsets previously used for calibrated light delivery, the see-through form factor preserves the wearer's view of the surroundings--relevant for clinical monitoring in supervised settings such as the intensive care unit. These results show that consumer XR glasses can serve as a dose-calibrated platform for wearable photostimulation using an open-source measurement chain, and provide groundwork for application-layer dose-response studies.

6
Optimizing Aqueous Humor Liquid Biopsy: Safety and Performance of a Short, Low-Dead-Space Ophthalmic Needle for Anterior Chamber Paracentesis

Singh, A. M.; Yeh, T.-C.; DeBoer, C.; Al-Moujahed, A.; Lin, J. B.; Smith, S. J.; Sanislo, S.; Janjua, K. A.; Lin, T.-C.; Almeida, D. R. P.; Mruthyunjaya, P.; Mahajan, V. B.

2026-09-02 ophthalmology 10.64898/2026.08.26.26361364 medRxiv
Top 0.4%
1.2%
Show abstract

Purpose: To evaluate the safety, procedural performance, sample recovery, and surgeon preference of an ophthalmic needle designed specifically for anterior chamber (AC) paracentesis. Methods: In this multicenter study, AC paracentesis was performed in clinic and operating-room settings using a 32-gauge x 4-mm needle with low dead space. The procedure was evaluated using a standardized physician survey. Prespecified outcomes included procedure-related adverse events (primary outcome), needle entry and handling, aspiration and sample recovery, comparative performance versus a 30-gauge needle, and physician preference for future use. Results: A total of 110 needle uses by eight surgeons were included. No ocular complications occurred, including lens or iris injury, hyphema, AC collapse, wound leak, hypotony, infection, or retinal complication, and no procedure required needle exchange or conversion to another device. Two technical events without ocular sequelae were noted, in which needle entry was partial thickness and did not reach the AC (1.8%; exact 95% CI, 0.2%-6.4%). Physicians rated needle entry, handling and sample recovery as good or excellent. Compared with a 30-gauge needle, the study needle was rated as at least comparable across all assessed domains. All surgeons rated it better or much better for intra-procedural safety and preferred it for future AC taps. Conclusions and Relevance: This short, 32-gauge low-dead-space ophthalmic needle demonstrated a favorable safety profile and was preferred over a 30-gauge needle by all surgeons. As aqueous humor liquid biopsy expands in clinical diagnostics and trials, an ophthalmic-specific needle design may help improve the consistency and safety of aqueous humor collection for molecular analysis and broader clinical use. Keywords: Anterior chamber paracentesis; Aqueous humor; Liquid biopsy; Low dead space; Ophthalmic needle

7
Dose-finding, experimental medicine evaluation of sodium valproate for the prevention of post-cardiac surgery myocardial injury

Roman, M.; Beasley, N.; Ladak, S. S.; Solomon, C. U.; Liao, W.; Lai, F.; Joel-David, L.; Aujla, H.; Condorelli, G.; Wozniak, M. J.; Codd, V.; Webb, T. R.; Brookes, C.; Murphy, G. J.

2026-09-02 cardiovascular medicine 10.64898/2026.08.30.26361746 medRxiv
Top 0.4%
1.1%
Show abstract

Background: A dose finding trial evaluated safety and adherence for pre-cardiac surgery administration of sodium valproate. Integrated multi-omics analyses of myocardium were used to characterise mechanisms underlying the treatment effects. Methods: Adults undergoing cardiac surgery were randomised 1:1:1:1 with concealed allocation to no treatment (Controls), sodium valproate 15mg/kg/day for 1-2 weeks, 15mg/kg/day for 4-6 weeks, or 25mg/kg/day for 4-6 weeks pre-surgery. The primary analysis evaluated adherence and toxicity. Myocardial injury was defined by high sensitivity serum troponin at 24 hours post-surgery. Single-nucleus Assay for Transposase-Accessible Chromatin with sequencing (snATACseq) and single nuclei RNA sequencing (snRNAseq) of myocardial biopsies collected at surgery assessed treatment effects on chromatin accessibility and gene expression. Candidate mechanisms were validated in in vitro. Results: The analysis cohort included 42 participants enrolled between January 2020 and August 2024. Non-compliance (38%) was highest with longer and higher dosing. Sodium valproate 15mg/kg/day for 1-2 weeks had the highest levels of complete treatment adherence (70%), with 20% experiencing moderate/severe drug related adverse effects. An as-treated analyses demonstrated reductions in troponin release in participants receiving Valproate[≤]14 days. Myocardial biopsies from trial participants demonstrated activation of hormetic p53 and Akt-GSK-3{beta} ferroptosis protection pathways. Treatment effects were not attributable to chromatin accessibility. Treatment >14 days resulted in a heart failure phenotype with suppression of ferroptosis protection pathways, endothelial mesenchymal transition, and increased myocardial injury. Conclusions: Sodium valproate 15mg/kg/day for [≤]14 days pre-surgery is well tolerated in adults awaiting cardiac surgery. This treatment was associated with upregulation of ferroptosis protection pathways and reductions in myocardial injury.

8
GLP-1/GIP Uptake, Indication, and Access Pathways Among US Adults in the Understanding America Study

Chaturvedi, R. R.; Gracner, T.; Perez-Arce, F.; Suen, S.-c.; Jin, J.; Orriens, B.; Pacula, R. L.; Sexton Ward, A.; Haile, R.; Kapteyn, A.

2026-09-02 endocrinology 10.64898/2026.08.28.26361368 medRxiv
Top 0.4%
1.1%
Show abstract

Importance: Evidence on GLP-1/GIP therapies is largely derived from trials enrolling selected populations or medical records that miss utilization outside healthcare channels. No nationally representative cohort has characterized real-world uptake, indications, and access. Objective: To characterize GLP-1/GIP prevalence, indication, clinical profile, and access. Design: Prospective cohort study with three GLP-1/GIP surveillance waves (March 2024, December 2024, October 2025). Setting: The Understanding America Study, an address-based, nationally representative panel of approximately 15,000 US adults aged 18+ years initiated in 2014. Participants: UAS participants responding to at least one surveillance wave (n=9150). Exposures: GLP-1/GIP use status (never vs any use, comprising current and former use), self-reported primary indication (diabetes, weight loss, or other), and access pathway (traditional vs non-traditional). Main Outcomes and Measures: Survey-weighted prevalence of GLP-1/GIP use, overall and by indication and access pathway; sociodemographic, cardiometabolic, treatment, and access characteristics; and smartwatch-derived resting heart rate, heart rate variability, maximum activity heart rate, step count, and sleep duration and variability. Results: Among n=9150 adults (1274 with any use; 60.9% female; median age 53 years), weighted prevalence increased 46%, from 8.2% (March 2024) to 12.0% (October 2025) representing 32 million. Weight-loss indications grew, reaching nearly half of use (4.1% to 5.6%); diabetes-indicated use was stable (5.3% to 5.4%). Users carried high cardiometabolic burden (obesity, 68.2%; diabetes, 53.6%) but diverged by indication: diabetes-indicated users were older (median, 59 vs 49 years), whereas weight-loss-indicated users were more often female (69.9% vs 51.3%) and healthier. One in three users (~9 million) had non-traditional access, especially in weight-loss-indicated users, of whom 33% had no conventional prescription; 41% used compounding, online, or foreign pharmacies; and, 43% lacked coverage. Non-traditional users were five times as likely to report an unlisted, likely compounded formulation (19.8% vs 4.1%). All p<0.05. Conclusions and Relevance: Real-world GLP-1/GIP use has grown rapidly and diversified substantially in indication, access, and population profile. One in 3 users obtained treatment through nontraditional channels largely invisible to claims data, raising long-term safety, efficacy, and coverage questions. GLIMMER provides a public, nationally representative longitudinal evidence base for future payer and provider decisions.

9
New tests for trials of very few patients using longitudinal data - a case-study in Autosomal Recessive Cerebellar Ataxias

Hendrickx, N.; Mentre, F.; Karlsson, M. O.; Hooker, A. C.; Traschütz, A.; Schüle, R.; PROSPAX Consortium, ; EVIDENCE-RND Consortium, ; Synofzik, M.; Comets, E.

2026-09-02 health informatics 10.64898/2026.08.28.26361588 medRxiv
Top 0.5%
1.1%
Show abstract

We propose two new tests to detect drug effects (DE) in trials of one to very few patients followed during two periods (before and after initiation of a treatment). Both methods use longitudinal natural history data to inform the estimation of each patient's DE. The first method uses a non linear mixed effect model (NLMEM) reflecting an expected natural history with a hypothetical drug effect, to estimate the Conditional Distribution of the Drug Effect (CDDE). The second method trains a Pareto Depth Analysis (PDA) algorithm, a machine learning based approach based on outlier detection, that we implement using data simulated under the NLMEM. We evaluated the two tests with a simulation study. We used data from the PROSPAX study in Autosomal Recessive Cerebellar Ataxias (ARCAs, to derive a NLMEM for the Scale for the Assessment and Rating of Ataxia score. The CDDE method provided controlled type I error and, in some scenarios, adequate corrected power, though sensitivity analyses showed vulnerability to misspecification. The PDA method demonstrated lower statistical power except with high score precision. These results highlight different strategies for quantifying treatment effects in ultra rare, patient' specific trials. They can inform methodological design for future ARCA precision therapies.

10
Adverse drug withdrawal event signals in FAERS and Eudravigilance databases: a stratified disproportionality analysis study

Khan, Z.; McCarthy, C.; Dalton, K.; Jungo, K. T.; Doherty, A. S.; Reeve, E.; Moriarty, F.

2026-08-31 pharmacology and therapeutics 10.64898/2026.08.29.26361707 medRxiv
Top 0.6%
1.0%
Show abstract

Background: Adverse drug withdrawal events (ADWEs) are a key safety concern during deprescribing but remain poorly explored in pharmacovigilance systems. Objectives: To identify and compare ADWE signals across drug classes, different drugs within drug classes, and across patient characteristics, countries, and over time. Methods: A case/non-case disproportionality analysis was conducted in FDA-FAERS and EMA-EudraVigilance pharmacovigilance databases, with stratification by age (adults: 18-64, older adults: [&ge;]65), sex (male/female), reporting time (2004-2023 in 5-year intervals), and country (for EMA data). Disproportionality analysis (quantitative signal detection) was used to detect signals between ADWEs and drugs using the proportional reporting rate (PRR[&ge;]2), reporting odds ratio (ROR>1), and information component (IC>0) with case count [&ge;]5. Results: Overall, 158,501 reports (FDA-FAERS 145,514; EMA-EudraVigilance 12,987) included drug-event pairs related to ADWEs. In FDA-FAERS, clobetasone (IC=5.58; PRR=79.18; ROR=176.90) showed the strongest ADWE signals, followed by hydromorphone (4.85; 29.94; 37.37), hydrocodone, and paroxetine. In EMA-EudraVigilance, ethyl loflazepate (IC=6.01; PRR=119.80; ROR=197.53), clobetasone (5.39; 102.73; 155.10), veralipride, and levomethadone had the strongest signals. Most drugs maintained positive ADWE signals in analysis stratified into adults and older adults. However, among the top 10 drugs (based on highest IC values), buprenorphine/naloxone, desvenlafaxine, and baclofen in FDA-FAERS (ICs 4.95-6.05) showed stronger signals in older adults. A sex-based difference was observed, with paroxetine, venlafaxine, and buprenorphine/naloxone showing a stronger positive signal in females in both databases, whereas several opioids had stronger signals in males versus females across both databases. Conclusion: This study suggests ADWE signals for some medications differ by age and sex, potentially indicating different risks for withdrawal effects.

11
The effect of high-dose glucocorticoids on opioid consumption in the first 24 hours after elective hip and knee arthroplasty: A natural experiment study of 47,317 surgeries in Eastern Denmark

Laigaard, J.; Moeller, M. O.; Olsen, M. H.; Overgaard, S.; Mathiesen, O.; Karlsen, A. P. H.

2026-09-02 pain medicine 10.64898/2026.08.31.26361793 medRxiv
Top 0.7%
0.7%
Show abstract

Background: In Denmark, perioperative high-dose glucocorticoid treatment were step-wisely implemented for total hip arthroplasty (THA), total knee arthroplasty (TKA), and unicompartmental knee arthroplasty (UKA). We aimed to estimate the effect of a single high dose of glucocorticoids on opioid consumption following primary THA, TKA, and UKA. Methods: This was a prespecified analysis of a multicenter natural experiment using electronic health record data. We included elective THA, TKA, or UKA surgeries performed in Eastern Denmark from 2017-2025. At each center, surgeries before implementation of high-dose glucocorticoids served as controls, whereas surgeries after implementation comprised the intervention group. The primary outcome was the between-group difference in cumulative 0-24h opioid consumption, which included preemptive end-of-surgery doses. The predefined minimal important difference was set at 5 mg IV morphine equivalents. Secondary outcomes were maximum 0-10 numerical rating scale (NRS) pain score and incidence of opioid-related adverse events within 24 hours, hospital length of stay, and days alive and out of hospital at 30 days. Results: A total of 47,317 surgeries performed at nine centers were analyzed: 13,010 controls and 34,307 in the intervention group. During the study period, five centers implemented high-dose glucocorticoids for THA patients, two for TKA/UKA patients. High-dose glucocorticoids were administered to 6% of patients before implementation versus 92% after. High-dose glucocorticoids resulted in a mean reduction of 3.8 mg intravenous (IV) morphine equivalents (95% CI 3.3;4.3). The intervention also reduced the maximum 0-24h NRS pain score by 0.8 points (99% CI 0.7;0.9), but there was no difference in adverse events, length of stay, or days alive and out of hospital. Conclusions: Implementation of high-dose glucocorticoids reduced 0-24-hour opioid consumption by 3.8 mg IV morphine equivalents after elective hip and knee arthroplasty. This difference was below the prespecified minimal important difference threshold. Online registration: https://doi.org/10.1101/2025.11.11.25339982

12
Individual-Level Counterfactual Analysis of SGLT2 Inhibitors Versus DPP4 Inhibitors in Diabetic Kidney Disease Using Causal Machine Learning

Yano, Y.; Nagasu, H.; Hiroshi, K.; Ohashi, M.; Isaka, Y.; Okada, H.; Nangaku, M.; Kashihara, N.

2026-09-03 health informatics 10.64898/2026.08.30.26361750 medRxiv
Top 1.0%
0.5%
Show abstract

Background: Traditional real-world studies comparing SGLT2 and DPP4 inhibitors on renal outcomes rely on propensity score matching, which causes high-dimensional data loss. We used causal machine learning (Causal ML) to unmask heterogeneous treatment effects in diabetic kidney disease (DKD). Methods: Using data from 4,588 patients within the Japanese J-CKD-DB-Ex registry, we implemented a doubly robust (DR) learning framework (Linear DR-learner with XGBoost) to compare SGLT2 and DPP4 inhibitors. Outcomes included the chronic eGFR slope and a composite renal endpoint ([&ge;] 50% eGFR decline or end-stage kidney disease). Heterogeneity was explored via causal SHAP and decision trees. Results: At the population level, SGLT2 inhibitors modestly slowed chronic eGFR decline (average treatment effect [ATE] = 0.14 [95% CI: -0.86, 1.15] mL/min/1.73m^2/year) and reduced composite endpoint risk by 9% (ATE: -0.09 [-0.11, -0.08]) versus DPP4 inhibitors. However, individual-level counterfactual analysis suggested that for the chronic eGFR slope, non-glinide users with stable pre-treatment trajectories who were also taking ACE inhibitors had a greater benefit from SGLT2 inhibitors (ATE: 2.95 [-0.68, 6.58]). Conversely, glinide users with steep pre-treatment decline had a greater benefit from DPP4 inhibitors (ATE: -8.98 [-16.11, -1.85]). For composite renal events, SGLT2 inhibitors had a 28% absolute risk reduction within the algorithmically identified high-risk subgroup (eGFR [&le;] 28.1 mL/min/1.73 m^2 and positive proteinuria; ATE: -0.28 [-0.33, -0.23]). Even non-proteinuric decliners demonstrated a 8% risk reduction with SGLT2 inhibitors (ATE: -0.08 [-0.10, -0.06]). Conclusion: Causal ML advances precision medicine in DKD, shifting from uniform prescribing to individualized, data-driven therapy targeting distinct intrarenal pathways.

13
Artificial Scientific Intelligence for Measurement-burden-aware Modelling and Interpretation of Multi-site Bone Mineral Density

Xiang, S.; He, H.; Xie, Z.; Cheng, C.-Y.; Li, H.; Liu, D.

2026-09-01 health informatics 10.64898/2026.08.30.26361665 medRxiv
Top 1%
0.4%
Show abstract

Agentic workflows can coordinate modelling, but balancing predictive performance, measurement burden and reproducibility is unclear. We developed DXA Agent, an agentic workflow for dual-energy X-ray absorptiometry (DXA) outcomes integrating planning, feature-model refinement, tools, provenance and hypothesis-generating interpretation. Models were independently developed and tested in UK Biobank (5,318 participants) and the National Health and Nutrition Examination Survey (NHANES; 3,777 participants), using cost-efficient and no-limit strategies. Across 20 UK Biobank and three NHANES bone mineral density sites, cost-efficient models achieved lower RMSE and higher R2 than the best conventional comparator, with median relative RMSE reductions of 10.9% and 9.9%, respectively. Classification was task dependent: UK Biobank osteoporosis averaged AUROC 0.839 and PR-AUC 0.182, whereas NHANES performance was comparable with conventional models. Higher-burden features did not consistently improve prediction. These retrospective, cohort-internal findings position DXA Agent as an inspectable, measurement-burden-aware research workflow requiring independent prospective validation.

14
Electronic health data exploring cardiorespiratory responses of transfusions in preterm infants: An international multicenter cohort study

Honore, A.; Rech, T.; Scrivens, A.; Binotto, I.; Zandvoort, C. S.; van der Staaij, H.; Peck, M.; Zivanovic, S.; Stanworth, S. J.; Hartley, C.; Dame, C.; Deschmann, E.; the Neonatal Transfusion Network,

2026-09-03 pediatrics 10.64898/2026.09.01.26361418 medRxiv
Top 1%
0.3%
Show abstract

Background and Objectives: Preterm infants are commonly transfused, yet direct cardiorespiratory effects of red blood cell (RBC) transfusions remain poorly understood. We explored the feasibility of using multicentre electronic health data (EHD) to study such cardiorespiratory responses. Methods: Highly granular routine EHD were collected from preterm infants born <32 weeks gestational age at three European centres. Heart rate, oxygen saturation, and respiratory rate were evaluated 12 hours before and after the RBC transfusion. Results: A total of 321 transfusions in 164 infants were analysed. Overall, there was no significant change in the rate of bradycardia and apnoea following transfusion. Cardiorespiratory parameters varied substantially between infants; e.g. 20% of transfusions were associated with an unexpected, significant increase in heart rate. Respiratory rate and oxygen saturation exhibited similarly heterogenous patterns following transfusion. In sub-group analysis, the proportion of transfusions with increased heart rate was significantly higher within the first two weeks than later (32% vs 13%, p=0.0019). Conclusions: Multicentre EHD extraction allows to identify otherwise masked short-term effects of RBC transfusions on cardiorespiratory parameters, possibly indicating cardiac or pulmonary overload. Such effects may vary with adaptation to anaemia. Analysing EHD may ultimately enable personalized transfusion practice.

15
Immortal time bias reproduces the reported survival benefit of conversion surgery in stage IV gastric cancer: a simulation study

Sah, B. K.; Li, C.; Li, J.; Zhu, Z.

2026-09-03 gastroenterology 10.64898/2026.09.01.26361986 medRxiv
Top 2%
0.3%
Show abstract

Background Conversion surgery for stage IV gastric cancer is supported by a pooled overall survival hazard ratio of 0.36 (95% confidence interval 0.32-0.40) and, in the largest international cohort, median survival of 36.7 versus 12.5-13.8 months on chemotherapy. Survival is measured from diagnosis; the median diagnosis-to-gastrectomy interval is 124 days, which patients must survive to be counted surgical. Methods We simulated cohorts of 3,177 stage IV gastric cancer patients from published parameters: background median survival 14.5 months; median diagnosis-to-surgery interval 124 days (category-specific 92-174 days). Surgery had no effect (true hazard ratio 1.00 by construction). Data were analysed as the literature analyses them (exposure fixed at baseline, follow-up from diagnosis), and by time-varying Cox and landmark analysis. Confounding by indication was added in a second scenario. Results Under immortal time bias alone the naive analysis returned a hazard ratio of 0.794 (95% simulation interval 0.743-0.851), median survival 16.8 versus 12.8 months. Time-varying Cox recovered 1.000 and landmark analysis 1.000-1.004. Bias scaled with the interval: 0.849 at 92 days, 0.715 at 174 days. Adding confounding, the naive estimate fell to 0.601 (0.560-0.644) at strength 0.5 and 0.356 (0.323-0.385) at strength 1.5, overlapping the published estimate; median survival 21.9 versus 8.7 months. Correcting immortal time alone left residual bias (hazard ratio 0.439). Conclusions The reported survival advantage of conversion surgery is reproducible where the operation does nothing; published estimates cannot distinguish benefit from bias. Resolving this requires individual patient data analysed with methods that assign person-time correctly, or completion of JCOG2301.

16
Clinical Reference Percentiles for AI-derived Epicardial Adipose Tissue: A Multicenter Study

Kamagate, A.; Shanbhag, A.; Buchwald, M.; Miller, R. J. H.; Khanna, S.; Zuhair Kassem, T.; Kwiecinski, J.; Bullock-Palmer, R.; Zhang, W.; Marcinkiewicz, A. M.; Yi, J.; Ramirez, G.; Lemley, M.; Killekar, A.; Kavanagh, P. B.; Liang, J. X.; Slipczuk, L.; Travin, M. I.; Alexanderson, E.; Carvajal-Juarez, I.; Packard, R. R.; Al-Mallah, M.; Ruddy, T. D.; deKemp, R. A.; Buechel, R. R.; Einstein, A. J.; Acampa, W.; Knight, S.; Le, V. T.; Mason, S.; Rosamond, T. L.; Miller, E. J.; Chareonthaitawee, P.; Berman, D. S.; Dey, D.; Di Carli, M. F.; Slomka, P.

2026-08-31 cardiovascular medicine 10.64898/2026.08.28.26360111 medRxiv
Top 2%
0.3%
Show abstract

Background and Aims: Epicardial adipose tissue (EAT) has emerged as an important cardiovascular biomarker that reflects both inflammatory and cardiometabolic risk. EAT volume and density vary significantly across populations, yet there is a lack of multicenter studies investigating the predictive value of population-specific EAT percentiles. Methods: In this multicenter study, we retrospectively analyzed low-dose computed tomography correction scans from 42,842 patients undergoing myocardial perfusion imaging. A derivation cohort of 15,082 patients was used to establish sex- and age-specific nomograms for EAT density and EAT volume indexed to body surface area. Percentile-based thresholds were tested for outcome prediction in a validation cohort of 27,760 patients. For clinical implementation, we developed an online EAT percentile calculator. Results: Percentile curves demonstrated increased BSA-indexed EAT volume and decreasing EAT density with age. Over a median follow-up of 3.6 years (IQR: 1.83 - 5.14), 4,956 patients experienced a nonfatal myocardial infarction or death. In multivariable Cox models, patients above the 95th sex- and age-specific percentile had significantly worse outcomes for BSA- indexed EAT volume [adjusted hazard ratio 1.30, 95% CI: 1.14 - 1.49, p < 0.001] and EAT density [adjusted hazard ratio 1.7, 95% CI: 1.51 - 1.92, p<0.001] when compared to patients below the 50th percentile (p<0.001). Conclusion: Age- and sex-specific EAT percentiles provide a clinically interpretable framework for contextualizing automated EAT measurements and identifying patients at increased cardiovascular risk. EAT density was a stronger prognostic marker and identified elevated risk even among patients with normal BMI, supporting its potential to provide information beyond conventional anthropometric assessment.

17
Immune Checkpoint Blockade Modifies Drug-Associated Toxicity Across Phenotypes and Time

Mukherjee, E. M.; Asiaee, A.; Park, D.; Krantz, M. S.; Stone, C. A.; Martin-Pozo, M.; Phillips, E. J.

2026-09-02 dermatology 10.64898/2026.08.31.26361880 medRxiv
Top 2%
0.2%
Show abstract

Importance: Immune checkpoint inhibitors (ICIs) produce diverse immune toxicities, but whether checkpoint blockade also modifies associations between other drugs and adverse events is poorly understood. Objective: To define ICI-associated toxicity organization and determine whether drug-associated adverse events and onset vary with ICI exposure and checkpoint pathway. Design and Setting: Cross-sectional analysis of deduplicated FAERS reports from 2016 through 2025; analyses performed in 2026. Participants: Among 13,701,106 deduplicated reports, 2,365,269 were cancer associated and 256,940 contained an ICI. Median age among cancer reports with observed age was 66 years (IQR, 56-75 years); 1,031,999 (43.6%) were female and 1,003,154 (42.4%) were male. Exposures: ICI exposure in any reported drug role, individual primary-suspect drugs, and checkpoint-pathway exposure. Main Outcomes and Measures: Reporting odds ratios (ORs), cross-organ adverse-event communities, adjusted primary-suspect drug x ICI interaction ORs for Stevens-Johnson syndrome/toxic epidermal necrolysis (SJS/TEN), drug reaction with eosinophilia and systemic symptoms (DRESS), acute generalized exanthematous pustulosis (AGEP), interstitial nephritis, drug-induced liver injury (DILI), and vomiting (VOM), and accelerated failure-time model time ratios for documented onset. Results: Of 3001 eligible Preferred Terms in cancer-associated reports, 2091 differed at a false discovery rate (FDR) less than .05. Four cross-organ toxicity communities were identified. Of 138 eligible drug-phenotype pairs, 65 had FDR-significant interactions, including moxifloxacin-SJS/TEN amplification (interaction OR, 101.72; 95% CI, 39.11-264.55), enfortumab vedotin-SJS/TEN attenuation (interaction OR, 0.17; 95% CI, 0.13-0.23), and omeprazole-interstitial nephritis amplification (interaction OR, 10.35; 95% CI, 7.62-14.05). Among 60,324 reports contributing to temporal analyses, ICI exposure was associated with longer adjusted documented time to onset for 5 of 6 phenotypes (time ratios, 1.37-1.59) but not AGEP (time ratio, 0.99; 95% CI, 0.67-1.46). Temporal associations also differed across checkpoint pathways. Conclusions and Relevance: ICIs were associated with a structured cross-organ toxicity landscape, phenotype-specific modification of drug-associated adverse events, and distinct temporal patterns across checkpoint pathways. These findings support checkpoint blockade as a modifier of drug-associated toxicity and motivate longitudinal and mechanistic validation.

18
Burden of fatigue in compensated chronic liver disease: findings from the multinational a:GAP Study

Choudhuri, G.; Akhundova-Unadkat, G.; Naidoo, N.; Morales-Castillo, M.; Guillaume, X.; Duijnhoven, R. G.; Safaei, A.; Swain, M. G.

2026-09-02 gastroenterology 10.64898/2026.08.28.26361618 medRxiv
Top 2%
0.2%
Show abstract

Background & Aims: Fatigue is a central symptom of chronic liver disease (CLD), substantially impacting health-related quality of life (HRQoL). This study aimed to further understand CLD symptomatology, including fatigue, and its impact on HRQoL from a patient perspective. Methods: Abbott Global Assessment of Patients unmet needs (aGAP) was a multinational, cross-sectional survey in adults with compensated CLD in China, India and Mexico, conducted between July and November 2024. Adult participants who self-reported that they had physician-diagnosed CLD and were experiencing fatigue completed a quantitative survey to assess symptom burden and included three HRQoL patient-reported outcome (PRO) questionnaires (Patient-Reported Outcomes Measurement Information System [PROMIS]-29+2, Work Productivity and Activity Impairment - Specific Health Problem version 2.0 [WPAI: SHP], Multidimensional Fatigue Inventory [MFI]). Results: Overall, 505 participants (China: 200; Mexico: 105; India: 200) completed the study. Participants reported that their CLD-related fatigue sometimes, often or always affected their self-esteem/confidence (45.1%) and ability to maintain or acquire new employment (38.6%). Most participants reported moderate (51.3%) or serious (26.9%) fatigue, with 33.5% experiencing fatigue every day or almost every day. Many participants felt their social life was negatively impacted by their fatigue (47.3%) and that there were related financial difficulties (53.9%). Use of validated PRO tools demonstrated severe fatigue (MFI: overall mean [SD] 13.9 [3.4] general fatigue and 13.4 [3.6] physical fatigue) as well as substantial levels of work and activity impairment (WPAI: SHP overall mean [SD] 53.0 [26.4]) and high levels of anxiety, pain interference, depression and sleep interference (PROMIS T-scores [&ge;]54). Conclusions: Fatigue has a substantial impact on HRQoL among adults with CLD across several countries, highlighting a global unmet need for targeted interventions to effectively identify and manage the condition.

19
Pain state-dependent multimodal assessment in non-specific low back pain: a pseudorandomized study design with implementation insights

Chozas Barrientos, B.; Hau, M.; Sirucek, L.; Langenfeld, A.; Wehrli, M.; Wirth, B.; Zoelch, N.; Devan, J.; Dudli, S.; Schweinhardt, P.

2026-08-31 pain medicine 10.64898/2026.08.26.26359760 medRxiv
Top 2%
0.2%
Show abstract

Background: Fluctuations in pain intensity are intrinsic to non-specific chronic low back pain (nsCLBP). Nevertheless, pain fluctuations have rarely been considered when investigating pathophysiological mechanisms. Therefore, a novel study protocol was developed and implemented to systematically assess the impact of fluctuating pain states on pain-related measures. Methods: The final study cohort consisted of 45 nsCLBP patients and 47 age- and sex-matched healthy controls (HCs). Patients participated in three visits, conducted during different pain states (i.e. clinically relevant pain, low-intensity clinical pain / pain-free, clinically irrelevant pain induced using a Qutenza 8% capsaicin patch). Pain fluctuations were monitored through online assessments every four days and guided the pseudorandomized visit scheduling. HCs participated in a single visit. Each study visit comprised a multimodal battery of pain-related measures. Results: 93.33% of patients completed all three visit types in a pseudorandomized order (chi-squared=1.50, p=0.826). Visit scheduling was possible due to the high self-report adherence (median=93.48%), unrelated to self-report burden (rho=-0.097, p=0.53). Study visits were conducted during different pain states, as indicated by: i) the significantly higher low back pain intensity in the clinically relevant pain visit (mean[SD]: 3.98[0.90]), compared to the low-intensity clinical pain (1.03[0.86]) and clinically irrelevant pain (1.13[0.82]) visits (p-values<0.001), as well as by ii) the successful induction of a moderate-to-high clinically irrelevant pain across assessments. Conclusion: Despite scheduling complexity and pain state transition uncertainty, a pain state-dependent pseudorandomized study design is feasible and could improve the understanding of nsCLBP mechanisms.

20
Cost-Aware Active Feature Acquisition for Differential Diagnosis under Realistic Clinical Availability Constraints

Bingham, J. C.; Arussy, N.

2026-08-31 health informatics 10.64898/2026.08.30.26361745 medRxiv
Top 3%
0.1%
Show abstract

Active Feature Acquisition (AFA) adaptively selects which diagnostic test to order next and offers a route to reduce unnecessary laboratory testing in acute care. Existing clinical AFA evaluations, however, assume every feature can be retrieved on demand and split data at the visit level, both of which inflate apparent performance. We re-evaluate cost-aware AFA under constraints designed to reflect deployment. From MIMIC-IV we constructed a cohort of 64,766 acute admissions (39,884 patients; 21 conditions; 55 features in 30 test panels) with a patient-level split, a 12-hour decision cutoff, and a per-patient availability mask from what was actually measured, and priced panels using the 2026 Medicare fee schedule under panel-level billing. We evaluated EIG-Cost, which scores each panel by Monte-Carlo Expected Information Gain penalised by its dollar cost, against eight published methods across budgets \30--$60 over five patient-level resamples. At a $30 budget, EIG-Cost achieved the highest macro-F1 (0.188, 95% CI [0.185, 0.191]) at the lowest cost ($17.28), exceeding the strongest baseline in all five resamples (p<0.001; Cohen's d=4.0), and led at every budget. Three of the eight methods collapsed to a vitals-only baseline (macro-F1 approx 0.040), acquiring nothing even at higher budgets, a genuine failure to adapt to availability rather than a budget limitation. Despite modest absolute accuracy, EIG-Cost's probabilities were well-calibrated (expected calibration error $0.048$). Under realistic availability constraints, clinical AFA is substantially harder than full-availability benchmarks imply, several published methods fail outright, and cost-aware information-gain scoring is a robust choice in this harder setting.